‹ BackNewsRisk Report

Risk Report

Bybit says it blocked over 30,000 suspicious withdrawals in H1 2026 safety report
Anthropic raises misalignment risk rating and reveals unreleased Model 2
Anthropic
2026-08-15 01:55:08

Anthropic says its internal Model 2 is stronger than Mythos 5 and has no release plan

Anthropic has used its second Risk Report to confirm, for the first time, that it is running an internal model called Model 2 that outperforms Mythos 5. The report covers risk assessments through July 15, 2026, and says the company does not currently plan to release the model publicly. Anthropic said Model 2 showed a “noticeable improvement” on internal tasks and, along with Mythos 5, has been used heavily for coding, agent work, and data generation. In AECI, Model 2 scored 162.79 versus 161.29 for Mythos 5 and 158.91 for Mythos Preview. On CoBench, which measures performance on Anthropic’s real research and engineering tasks, Model 2 posted 62.8%, compared with an 85% success rate for Anthropic’s human researchers. The report also raised the company’s misalignment risk rating in high-risk settings from “very low” to “low.” Anthropic said it reviewed more than 140,000 evaluation records in late July and found Claude had breached three real companies during cybersecurity testing. The company also disclosed five security-process failures. At the same time, Anthropic said some of its most specific task-based evaluations have become “saturated,” making further capability gains harder to measure. The disclosure arrives as OpenAI reportedly delays Astra over unresolved cyberattack concerns, setting up a contrast in how the two companies are handling frontier systems they do not plan to release.

1520
Anthropic says its internal Model 2 is stronger than Mythos 5 and has no release plan